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We investigate the (non)-existence of universal automata for some classes of automata, such as finite 
automata and pushdown automata, and in particular the influence of the representation and encoding 
function. An alternative approach, using transition systems, is presented too. 

1 Introduction 

It is well known that there exist universal Turing machines (UTM). Such a UTM simulates any special 
Turing machine (TM) M in a certain way. There are several ways of simulation. One is that a UTM U 
simulating a TM M with input w halts if and only if M halts on input w. Another possibility is that any 
computation step of M is simulated by U using some number of steps which are be restricted by some 
complexity function. In very small UTM's this can be exponential. 

Almost all UTM's consuucted so far are deterministic, simulating deterministic TM's. In (4[ it has 
been shown that there exist UTM's simulating all special TM's with complexity constraints. These 
complexity constraints, for space or time, are from a subclass of all primitive recursive functions over 
one variable. The UTM's have the same complexity constraints. 

In both cases, general TM's and those with complexity constraint, the specific TM M and its input 
w € £(M)*, where £(M) is the alphabet of M, have to be encoded. Such an encoding, and also the 
decoding, can be achieved by deterministic finite state transducers (DFST), which means that encoding 
and decoding is bijective. The input for a UTM U, to simulate M with input w, can then have the form 
c m (M)#Ci(w) where c m , C{ are the encoding functions for M, w, respectively. 

If one intends to construct universal machines for weaker automata classes it should be kept in mind 
that encoding and decoding for such automata should not exceed the power of deterministic versions of 
those machines. Otherwise too much power and information could be hidden in the encoding. 

In Q it has been shown, under this condition, that there don't exist universal 1-way finite automata 
(FA), neither deterministic (DFA) nor nondeterministic (NFA) ones. The proof uses arguments on the 
number of states of such automata. 

So the question arises whether there exist universal universal pushdown automata (UPDA), and if so 
if encoding and decoding can be achieved by DFST's, or if deterministic pushdown transducers (DPDT) 
are necessary. For general TM's DFST's suffice for encoding and decoding (e.g. 1151"). 

2 Transducers 

Pushdown uansducers (see e.g. fflEl) are just the analogon to finite state transducers, i.e. pushdown 
automata with output. 

Formally: 

A (non-deterministic) pushdown transducer (PDT) is a construct (Q,T.j,'E, ,A,$,qo,Qf,p) where 
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Q set of states 

Z; input alphabet 

Z output alphabet 

A pushdown alphabet 

$ G A stack bottom symbol 

q G Q initial state 

Qf CQ set of final states 



p C Q x Zj* x Z* x A* x A* x Q the transition relation . 



A deterministic pushdown transducer (DPDT) is the deterministic version of PDT. That is, p is a 
function p : Q x Z* x A* -> Z* x £. 

In the sequel a normal form of PDT's will be considered, being quasi lettering in input and pushdown, 
i.e. p CQx (Z ; -U{A» x (Z U{A}) x (Au{A}) x (Au{A}) x Q. 

Contrary to regular languages (REG) which are closed under finite state transductions the context- 
free languages (CF) are not closed under pushdown transductions, neither non-deterministic nor deter- 
ministic. Even linear context-free languages give sets outside CF if a deterministic pushdown transduc- 
tion is applied. This is shown by the following examples of DPDT's where T denotes the function defined 
by a deterministic transducer. 

Let Li = {0 n l \n>0}£ REG and 

T\ = ({90,91}, {0,1}, {0,1}, {90}, {91}, Pi) with 
Pi = {(go,0,0,$,0$,(7o), (^o,0,0,0,00,^o), (?o,l, A, $,$,42), 
(90, 1, A, 0,0,91), (91, A, 1,0, A, 91), (91, A, A, $,$,92)}- 
Then t(Li) = {0"l n | n > 0} REG. 

Let L 2 = {0"1"0 I n > 0} e LIN and 

?2 = ({90,91^2,93}, {0,1}, {0,1}, {q }, {93}, P2) with 

p 2 = {(90,0, 0,$,0$,9o), (9o,0,0,0,00,9o), (90, 1, 1,0, 10,90, 

(91, 1,1, l,U,9i), (#i,0, A, 1,1, 92), (92, A, 0,1, A, 92), 

(92, A, 1,0, A, 92), (92, A, A, $,$,93)}. 
Then t(L 2 ) = {0"l n 0"l n | « > 0} CF. 

Let L 3 = {wcw R c I w € {0, 1}+} G LIN and 

T3 = ({90, 9i, #2, 93}, {0,1}, {0,1}, {90}, {93}, P3) with 

P3 = {(9o,0,0,$,0$,9o),(9o,l,l,$,l$,9o),(9o,0,0,0,00,9o), 

(9o,0,0, 1,01, 90), (90, 1,1, 0,10,9o), (90, 1,1, 1,11, 90), 
(90, c,c, 0,0,9i), (90, c,c, 1,1, 91), (91, 0,0,$, 0$,9i), 
(9i,l,l,$,l$,9i),(9i,0,0,0,00,9i),(9i,0,0,l,01,9i), 
(91, 1, 1, 0, 10,9i), (91, 1,1, 1,1 1,91), (91, c,c, 0,0,92), 
(91, c,c, 1,1, 92), (92, A, A, 0,0,92), (92, A, A, 1,1, 92), 
(92,A,A,c,c,9 2 ),(92,A,A,$,$,9 3 )}. 
Then t(L 3 ) = {wc/cwcw s c [ w G {0, 1}+} CF. 

One also might consider 2-way FST's (2FST). However, REG is not closed under 2-way finite state 
transductions, as can be seen from the following example of a 2DFST (L,R,M denote move to left, right, 
or not, respectively). 

Let U = {awb \ w G {0, 1}*} G REG and 
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Ta = ({?o,?i,?2,93},{0,l,a,6},{0,l,a},{9o},{?4},P2) with 
Pa = {(qo,a,a,R,qo),(qo,0,0,R,q ),(qoAA,R,qo),(qo,b,X,L,q 1 ), 
(q i ,0,X,L,q l ),(qi,\,X,L,q l ),(qi,a,a,R,q 2 ), 
(q 2 ,0,0,R,q 2 ), (92, 1, 1,^,92), (?2,M,Af,g 3 ). 
Then t(L 4 ) = {awaw | w G {0, 1}*} CF. 



3 Representations 

A representation of FA, PDA, TM's etc. has to contain information on the set of states, initial and final 
states, alphabets, and the set of transitions: 

(<2,E,A,<2o,<2/,#) with e.g. RCQxE'xA'xA* x Q for a PDA. 
Usually R is represented by an ordered list of elements from R, together with lists for <2o and Q f. For 
non-deterministic machines one might also allow repetitions of list elements. This can give a regular set 
of representations if Q, E, and A are fixed. 

FA and PDA usually are represented by the list R of their transitions, putting together the tuples 
(q,x,q') or (q,x,y,y' ,q') for FA or PDA respectively, where xG (ZU{A}), andy,/ G (AU{A}). 

For FA one has a representation R G (Q ■ (EU {*}) • £>)+, and for PDA R G (Q ■ (EU {*}) • (AU {*}) • 
(AU {*}) • <2) + where * stands for A. Together with the input w this gives a representation R(M)#w. But 
one might think also of a representation w#R(M) or even R(M) lu#luh' where lu is the shuffle operation, 
and # a separation symbol. 

If the first version of representation is encoded by a DFST T the result is a word z(R(M))z(#)z(w) 
where z is the function associated to T. This follows from the fact that T is working 1-way. The 
same conditions hold for the 'inverse', namely for a DPDT T' with associated function z' . Furthermore, 
z'(z(R(M))z(#)z(w)) = R{M)#w has to hold. 

Considering representations of FA with arbitrary sets of states Q and arbitrary alphabets £, states qu 
can be represented by qa k (0 < k) and symbols x m by xa m (0 < m) over the finite alphabet {q,x,a}. Then 
the set of all representations of finite automata is given by 

F = (M • {a} + ■ ({x} ■ {a}+ U {*}) ■ {q} ■ {a}+)+. " 

Clearly, F G REG. But note that this holds only for non-deterministic FA. In the case of DFA there is 
the condition that a pair (qk,x m ) of state and symbol, represented by qa k and xa m , can appear only once 
as first components in the list of transitions. 

An analogous property also holds for PDA. 

Since REG and CF are closed under FST mappings, encoding (and decoding) will not lead out from 
these classes. However, the examples above show that this does not hold for PDT mappings. Therefore, 
not to gain too much power it might be reasonable to have a condition that those PDT mappings used for 
encoding and decoding are not leading out from the class CF. 

4 Universality? 

In the sequel we shall consider PDA M accepting languages L(M) C {0, 1}*. Assume that there exists 
a universal PDA U simulating all specific PDA M over {0, 1}. Denote this class by CF2. Let be the 
alphabet of U. Then, with the first version of representation, 
L(U) = {z{R(M))z{#)z(w) I L(M) G CF 2 ,w G L(M)} G CF. 
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Consider now the special regular (context-free) languages {w} C {0, 1}*. A representation of a DFA 
M, being also a PDA or 2DFA, accepting exactly the language {w},e.g. looks like/? (M) = qowqi = <p(w). 
Together with an input v G {0, 1}* the representation has the form R(M)#v = qowq\#v. A DFST maps 
this into t(<p(w))t(#)t(v). Clearly, 

S = {t(<p(w))t(#)t(v) |v,wG{0,1}*} 

= {T(<p(w)) I w G {0, 1}*} • {t(#)} • t({0, 1}*) G REG. 

Now 

L(U)nS= {t(<p(w))t#t(v) I w,v G {0, l}*,v = w} 

= {t(<p(w))t(#)t(w) I w G {0, 1}*} G CF 
since CF is closed under intersection with regular sets. 
Applying the 'inverse' DFST mapping %' yields 
z'(L(U)r\S) = {(p(w)#w I w G {0, 1}*} G CF. 

Another DFST mapping \j/ with \j/(qo) = Will) = ^> V(0) = 0' V(l) = 1 § ives 
VA(T'(L(f/)n5)) = | w G {0, 1}*} G CF, 

a contradiction. 

These considerations can be summarized as 

Theorem 1: If encoding and decoding of specific finite or pushdown automata have to be achieved 
by DFST then there doesn't exist a universal finite automaton, or 2-way finite automaton or pushdown 
automaton, simulating all specific finite automata. □ 

It should be remarked, however, that the proof of this theorem cannot be used to show that the 
statement also holds for all quasi lettering finite automata. The reason is that the DFST for encoding has 
to know the length of w for the representation qoxiqi ■ --x^qk where x,- G {0, 1} and all qi are different. 

Therefore we give another proof that this theorem also holds for quasi-lettering automata. For non- 
deterministic (quasi-lettering) FA it can be assumed that there is exactly one initial and exactly one final 
state, and that the first element in the list has the form (qi,x, q) where qi is the initial state and x G E, and 
that the last element has the form (^2,^,^2) where qi is the final state. 

Let = (<2, {0, , {qi},Rn) a FA, illustrated in Figure 1, where 

M = {i,i',i"JJ'J",k,k',k",£,£',£") and 
2 < i,i',i",j,j',j",k,k',k", £,£',£" < \Q\ are fixed and 
R{M ll ) = { {q\ ,0,qi),{qi,l,q k ), (q? ,0,q f ), (q t i , 1 , qe» ) , 

(q k »,0,qj"),{qk>,i,qe)i( ( ijAq2),(qe,i,q2),(q2A,q2)}- 




Figure 1: FA 

L(M JU ) = {000, 01 1 , 100, 111} implies i = i' = i" , j = f = j", k = k' = k" , £ = £' = £" . Note that Af f 
is also a PDA. 
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An encoding of R(M I1 ) is given e.g. by 

t(/?(M jU )) = qaxbqa' ■ qaxcqa k ■ qa'xbqa-i' ■ qa 1 " xcqa?" ■ 

qa k xbqaJ • qa k xcqa e • qaxbqa 2 • qa e xcqa 2 ■ qa 2 * qa 2 
T is a DFST mapping. Note that t actually depends on Q. 

Now assume that 

U = {z(R(M))z(#) |MeQLFA, {000, 01 1,100, 111} CL(M)} 
•t({000,011,100,111}) 

is context-free, where QLFA denotes the class of quasi-lettering FA with set of states Q, initial state 
qi, final state qi, and alphabet {0, 1}. 

Define S = {t(/?(M^))t(#) | 2 < i,i> ,i" ,j,f J" ,k,k' ,k" ,£,£' ,£" < \Q\} 

■t({000,011,100,111}). 
Obviously, S is regular, implying that U HS is context-free. But 
U n S = {qaxbqa' ■ qaxcq k ■ qa'xba f ' ■ q'xcqa f ' ■ qa k xbqa-i ■ qa k xcqa i ■ qa^xbqa 2 

■qa e xcqa 2 ■ qa 2 * qa 2 | 2 < i, j, k, i < | Q\ } 
is not context-free, a contradiction. 



If 2-way pushdown transducers are allowed however, and the encoding is not required to have the 
form t(R(M)t(#)t(w), e.g. in a more general form z(R(M)#w), it is possible to construct a universal 
PDA, following an idea by Gh. Paun et. al.. To be more precise, there exists a quasi-lettering universal 
PDA simulating all lettering PDA in this way. 

LetM= (Q,L,A,d,q ,{qf}) with Z = {0, 1} be a lettering PDA,and/?(M) C{QxIxAxAx Q}* 
be its representation. Suppose we allow z(R(M)#w) to be transduced by a 2-way PDT (2PDT). Then one 
can choose the coding x(R(M)#x\_x 2 . ..x„) = x\0 (R(M))x 2 o (R(M)) . . .x n %o{R(M)) where x t G £, and 
a(R(M)) is a coding of R(M) in a fixed alphabet. Note that a depends on Q. For simplicity let a = l at 
first where i is the identity function. Then x can be calculated by a 2PDT that works as follows: 

1. First go right until the place after # and print the first symbol of w. 

2. Go left and push a G A to the stack until reading #. 

3. Go left to the beginning of the input. 

4. Print R(M) until reading #. 

5. Go further right and simultaniously pop a G A from the stack. On empty stack check if the end of 
the input has been reached. 

(a) If so, print '%', go to the beginning of the input and copy R(M) one last time and halt. 

(b) Otherwise print the current input symbol and repeat from step 2. 

Consider the following PDA U (Figure 2) with Z v = {0, 1} U (Q x E x A x A x E) and A v = {a} U Q 
for all possible <2,Z, A. The Zj/, Ay are not finite, but when chosen an appropiate encoding a for the rep- 
resentation of simulated automata, the following idea for an UPDA works with finite alphabets Z[/,A[/. 
For now, assume any transition (q,x,a,b,q') in the simulated automaton to be an atomic symbol of our 
UPDA. 

At any time t/'s stack consists of a word qS m S m -i ...do where 5 = 8 m . . . do is the stack content and 
q the current state of M during a simulation. In state R, U first reads a symbol x ( from the input (word 
w) and afterwards checks whether the simulated PDA could have read x,- by travelling through the input 
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Figure 2: Universal PDA 



R{M) and looking for a transition (q,xi, 8 r , 8 w ,q') with q being the state M is currently in, and storing the 
new state and changed stack content of M in its own stack. 

In the construction we have used (q,x,a,b,q') etc. as one symbol, but it works also if an encoding 
a over a finite alphabet L v is used. E.g., o({qi,x,8j,8k,qe)) = DSA 1 xPB j PB h SA ( with x G £ = {0, 1} 
gives such an encoding. Then, point 4 in the transduction x would contain additional steps to encode a 
transition. E.g. to read symbol qi G Q and print Sa' and so on. Also, the UPDA U must have additional 
components for decoding in all states that have outgoing transitions reading a tuple (q,x,a,b,q') from 
the input. 

This can be achieved as follows. Let Ly = {A,B, D, P, S, T, 0, 1 } where D, P, S, T are markers. D is the 
deliminator of DSA'yPB J PB k SA e encoding (qt,y, 8j, 4, qi), and T the deliminator of a block a(R(M)). 

In a step U, after reading x G {0, 1}, goes into different states according to x. The stack of U contains 
SA m PB n PB r ■■■$ where $ is the bottom symbol. U non-determistically goes to some DSA i yPB j PB k SA e 
within a(R(M)). lfi = m ,x = y, and j = n then SA m PB n ■ ■ ■ $ in the stack is replaced by SA e PB k ■■■$. 
Otherwise, U goes into a sink. Note that U is a quasi-lettering PDA. 

With a slight modification in the encoding it can be shown that the result also holds for PDA M not 
being lettering for the pushdown alphabet A. 



5 Transition Systems 

In this section we present an alternative approach to investigate universal automata. 

In order to specify what we mean by 'universal PDA or 'universal grammar', we introduce the 
notion of Transition Systems, a generalisation of finite automata that models any kind of computational 
device having an internal state that can be altered by the occurrence of an action during the course of a 
computation. 



Definition 1: (Transition System) 

A Transition System is a quintuple (S,L, 8,Sq,Sf) with 
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S a set of states, 

£ an alphabet of transitions 

<5CSx(£U{A})x<2 a transition relation 

So C 5 a set of initial states 

Sp C S a set of final states 

We write s — -4 s' for (s,t,s') G 5. A transition system is called finite if SU T is finite. The transition 
relation in a transition system can be extended to finite sequences of transitions: 

A 

• s — > s for all states s and the empty sequence X . 

• s s' iff a state s" exists, such that s 5" A s" s'. 
Let denote the transitive and reflexive closure of — K 

The language of the transition system A is 

L(A) = {w G I*|5o j e ^o,*/ G 

Any finite automaton A = (Q,L, 8,qo, Qp) is also a transition system by definition. However, in a 
transition system Q and <2f are generally not finite. Any PDA A = (£?,£, A, 5,<7o> £2f) defines a transition 
system whose states are all possible configurations c G 2 x A* of the PDA, and transitions are defined by 

(q,wx) -A (?',wy) (q,a,x,y,q') G 5. 

Definition 2: (Universal Transition System) 

Let X C RE be a language class below RE. The TS A = d^So^f) is called X-universal iff for 
any language L G T* in X there is a state sl £ S such that L((S,'L,8,si,Sf)) = L. 

Any computing device that unambiguously defines a transition system will be considered X-universal 
if its transition system is. 

Obviously there is a REG-universal system, namely the disjoint union of all the possible NFA. The 
interesting question is however, whether or not such a REG-universal system can be defined by an NFA. 
The following lemma recovers a theorem from [3 ] from a new perspective. 

Lemma 1: There is no REG-universal NFA. 

Proof : For arbitrary n G N,a G L L n = {a n } G REG. Any REG-universal system U must have a state 
sl„ from that on exactly n steps can be made. If the set of states in U was finite, this state could not exist 
for n > \S\,$oU must have an infinite set of states. Since any NFA defines a finite transition system, no 
REG-universal system can be defined by an NFA, and therefore no REG-universal NFA exists. □ 

6 Outlook 

We have shown that there doesn't exist universal 1-way or 2-way finite automata nor pushdown automata 
if encoding and decoding have to be done by deterministic finite state transducers. However, if 2-way 
deterministic pushdown transducers are allowed and for encoding of a repetition of the specific pushdown 
automaton to be simulated, depending on the length of input, a universal pushdown automaton can be 
constructed. Further research has to be done on encoding and decoding device. Our conjecture is that 
1-way deterministic are not sufficient for the existence of a universal pushdown automaton. 
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